Papers with computational research

11 papers
Achieving Common Ground in Multi-modal Dialogue (2020.acl-tutorials)

Copied to clipboard

Challenge: tutorial focuses on three main topic areas: grounding in human-human communication, dialogue systems and multi-modal interactive systems.
Approach: This tutorial examines the use of computational dialogue research to design grounding modules and behaviors in cutting-edge systems.
Outcome: This tutorial examines the results of recent research on grounding in human-human communication . it shows how these results lead to rich and challenging opportunities for doing grounding more flexible and powerful ways .
Measuring and Modeling Language Change (N19-5)

Copied to clipboard

Challenge: This tutorial will help researchers answer questions fundamental to the social sciences and humanities .
Approach: This tutorial is designed to help researchers answer questions in the social sciences and humanities . it synthesizes recent computational techniques for handling and modeling temporal data .
Outcome: The tutorial will synthesize recent techniques for handling and modeling temporal data, such as dynamic word embeddings, and identify useful tools for social scientists and digital humanities scholars.
Dreaddit: A Reddit Dataset for Stress Analysis in Social Media (D19-62)

Copied to clipboard

Challenge: Existing computational studies on stress only focus on domains such as speech or Twitter . a corpus of social media text is used to identify stress .
Approach: They propose a text corpus of lengthy social media data for detecting stress . they use 190K posts from five different categories of Reddit communities .
Outcome: The proposed corpus of social media data can be used to identify stress . it includes 190K posts from five different categories of Reddit communities .
Rethinking Research on Stereotypes: An Analysis through Social Psychological and Computational Perspectives (2026.findings-acl)

Copied to clipboard

Challenge: Existing research on stereotypical biases ignores literature on them and results in resource wastage.
Approach: They argue that stereotypes are social constructs shaping human perception and behavior that can produce harmful outcomes under specific conditions.
Outcome: The proposed models can inherit and amplify stereotypes under certain conditions.
Blow the Dog Whistle: A Chinese Dataset for Cant Understanding with Common Sense and World Knowledge (2021.naacl-main)

Copied to clipboard

Challenge: Cant is important for understanding advertising, comedies and dogwhistle politics . currently, there are very few resources available for the research of cant .
Approach: They propose a large and diverse dataset for creating and understanding cant from a computational linguistics perspective.
Outcome: The proposed dataset can be used to test word embedding similarity and pretrained language models.
Do Models of Mental Health Based on Social Media Data Generalize? (2020.findings-emnlp)

Copied to clipboard

Challenge: Existing literature on the validity of proxy-based methods for annotating mental health status in social media has raised new concerns regarding their use in clinical applications.
Approach: They explore the generalization ability of machine learning classifiers trained to detect depression in individuals across multiple social media platforms.
Outcome: The proposed methods show that they can be used to train and analyze large datasets and that they are robust to large dataset sizes.
HiRAS: A Hierarchical Multi-Agent Framework for Paper-to-Code Generation and Execution (2026.findings-acl)

Copied to clipboard

Challenge: Existing approaches to automate computational research use fixed sequential agent pipelines with weak global coordination, which limits their robustness and overall performance.
Approach: They propose a hierarchical multi-agent framework for end-to-end paper reproduction that employs supervisory manager agents to coordinate specialised agents across fine-grained stages.
Outcome: The proposed framework improves the paper2code benchmark and significantly reduces hallucination in the evaluation.
Trends, Limitations and Open Challenges in Automatic Readability Assessment Research (2022.lrec-1)

Copied to clipboard

Challenge: Readability assessment is the task of evaluating the reading difficulty of a given piece of text.
Approach: They examine the common approaches used for automatic readability assessment and identify their shortcomings and some challenges for the future.
Outcome: The proposed models are compared with existing models and are based on existing ones.
A Survey of Cognitive Distortion Detection and Classification in NLP (2025.findings-emnlp)

Copied to clipboard

Challenge: despite momentum in natural language processing, the field remains fragmented . inconsistencies in CD taxonomies, task formulations and evaluation practices limit comparability .
Approach: This review provides a comprehensive review of 38 studies spanning two decades . they map how CDs have been implemented in computational research and evaluate the methods applied.
Outcome: The paper presents the first comprehensive review of 38 studies spanning two decades . it summarises common task setups and highlights persistent challenges to support more coherent research.
NORMSAGE: Multi-Lingual Multi-Cultural Norm Discovery from Conversations On-the-Fly (2023.emnlp-main)

Copied to clipboard

Challenge: Existing methods to understand acceptable behavior have focused on a single culture and manually built datasets from non-conversational settings.
Approach: They propose a framework to automatically extract culture-specific norms from multi-lingual conversations.
Outcome: The proposed framework extracts culture-specific norms from multi-lingual conversations.
Mapping the Podcast Ecosystem with the Structured Podcast Research Corpus (2025.acl-long)

Copied to clipboard

Challenge: a dataset of over 1.1M podcast transcripts is largely comprehensive of all English language podcasts available through public RSS feeds from May and June of 2020.
Approach: They propose to build a large-scale open dataset of podcast transcripts that includes metadata, speaker roles, audio features and speaker turns for a subset of 370K episodes.
Outcome: The proposed dataset is largely comprehensive of all English language podcasts available through public RSS feeds from May and June of 2020.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations